Repository navigation
[bench] sparse bitpacked filter and take - #9900
robert3005 wants to merge 2 commits into
Conversation
Merging this PR will improve performance by 34.45%
|
| Mode | Benchmark | BASE |
HEAD |
Efficiency | |
|---|---|---|---|---|---|
| ⚡ |
Simulation | density_sweep_dense_runs[0.001] |
48.2 µs | 30.1 µs | +60.13% |
| ⚡ | WallTime | bitpack_blocked_compress_avx2 |
7.6 µs | 6.7 µs | +12.89% |
| 🆕 | WallTime | filter_avx512[64] |
N/A | 2.2 µs | N/A |
| 🆕 | WallTime | filter_avx512[8] |
N/A | 1.1 µs | N/A |
| 🆕 | WallTime | filter_avx512[80] |
N/A | 2.3 µs | N/A |
| 🆕 | WallTime | threshold_avx512[64] |
N/A | 2.9 µs | N/A |
| 🆕 | WallTime | threshold_avx512[8] |
N/A | 1.8 µs | N/A |
| 🆕 | WallTime | threshold_avx512[80] |
N/A | 3.2 µs | N/A |
| 🆕 | WallTime | filter_avx2[64] |
N/A | 2.4 µs | N/A |
| 🆕 | WallTime | filter_avx2[8] |
N/A | 1.1 µs | N/A |
| 🆕 | WallTime | filter_avx2[80] |
N/A | 2.5 µs | N/A |
| 🆕 | WallTime | threshold_avx2[64] |
N/A | 3 µs | N/A |
| 🆕 | WallTime | threshold_avx2[8] |
N/A | 1.9 µs | N/A |
| 🆕 | WallTime | threshold_avx2[80] |
N/A | 3.3 µs | N/A |
| 🆕 | WallTime | filter_neon[64] |
N/A | 2.9 µs | N/A |
| 🆕 | WallTime | filter_neon[8] |
N/A | 1.1 µs | N/A |
| 🆕 | WallTime | filter_neon[80] |
N/A | 3.1 µs | N/A |
| 🆕 | WallTime | threshold_neon[64] |
N/A | 4.5 µs | N/A |
| 🆕 | WallTime | threshold_neon[8] |
N/A | 3.1 µs | N/A |
| 🆕 | WallTime | threshold_neon[80] |
N/A | 4.8 µs | N/A |
Tip
Curious why performance improved? Comment @codspeedbot explain why performance improved on this PR, or directly use the CodSpeed MCP with your agent.
Comparing rk/fastlanes-sparse-benchmarks (6cd3ac0) with develop (1ee4e3a)
Footnotes
-
359 benchmarks were skipped, so the baseline results were used instead. If they were deleted from the codebase, click here and archive them to remove them from the performance reports. ↩
|
This PR has been marked as stale because it has been open for 14 days with no activity. Please comment or remove the stale label if you wish to keep it active, otherwise it will be closed in 7 days |
Signed-off-by: Will Manning <will@willmanning.io> Signed-off-by: "Robert Kruszewski" <github@robertk.io>
f17daff to
6078880
Compare
|
Do we need 72 benchmarks? |
|
you're correct that we don't |
Adds sparse bitpacked filter and take benchmarks for
u8,u16,u32, andu64, providing a baseline for the extraction optimization in #9723.